Skip to content

Top 100 AI Safety Repositories

Back to Home

Ranking

RankingProject NameStarsForksLanguageOpen IssuesDescriptionLast Commit
1wasp18,7181,471TypeScript800The batteries-included full-stack framework for the AI era. Develop JS/TS web apps (React, Node.js, and Prisma) using declarative code that abstracts away complex full-stack features like auth, backgr...2026-09-03
2Zero10,7941,329TypeScript0Experience email the way you want with Mail0 – the first open source email app that puts your privacy and safety first. Join the discord: https://mail0.link/discord2026-05-26
3superagent6,727960TypeScript9Superagent protects your AI applications against prompt injections, data leaks, and harmful outputs. Embed safety directly into your app and prove compliance to your customers.2026-08-25
4Awesome-OSINT-List4,348583Shell1📡 Comprehensive collection of OSINT tools for cybersecurity professionals, researchers, and bug bounty hunters. Topics: information gathering, reverse search, red team, trust & safety, AI.2026-08-12
5goclaw3,5851,047Go144GoClaw - GoClaw is OpenClaw rebuilt in Go — with multi-tenant isolation, 5-layer security, and native concurrency. Deploy AI agent teams at scale without compromising on safety.2026-09-03
6sunnypilot2,0731,515Python17sunnypilot is an open source driver assistance system. sunnypilot offers the user a unique driving experience for over 350 supported car makes and models with modified behaviors of driving assist enga...2026-09-03
7valqore1,93677Python0Safety-first guardrails for AI-driven cloud and Kubernetes operations2026-09-04
8agentic_coding_flywheel_setup1,628184Shell5Bootstraps a fresh Ubuntu VPS into a complete multi-agent AI development environment in 30 minutes: coding agents, session management, safety tools, and coordination infrastructure2026-09-04
9Awesome-Jailbreak-on-LLMs1,612125-0Awesome-Jailbreak-on-LLMs is a collection of state-of-the-art, novel, exciting jailbreak methods on LLMs. It contains papers, codes, datasets, evaluations, and analyses.2026-08-10
10cc-safety-net1,52276TypeScript1A pre-execution guard for AI coding agents. It blocks destructive Git and file system commands, plus common attempts to access sensitive files, before a tool call runs. Supports Amp Code, Antigravity ...2026-09-04
11Internal-Safety-Collapse1,179196Python0We built an adversarial codespace setup. Place any AI agent into a normal workflow inside it, and the agent will fill in whatever is missing.2026-08-27
12shellfirm92733Rust1Safety guardrails for ai coding agents and human terminal commands2026-05-15
13Synthid-Bypass923129Python15Ai safety research showing a working bypass to Google's synthid on Nano Banana Pro2026-07-01
14AgentDoG69434Python1A Diagnostic Guardrail Framework for AI Agent Safety and Security2026-06-08
15polymarket-mcp-server666139Python6🤖 AI-Powered MCP Server for Polymarket - Enable Claude to trade prediction markets with 45 tools, real-time monitoring, and enterprise-grade safety features2026-07-30
16ai-scanner656110Ruby0AI model safety scanner built on NVIDIA garak2026-08-31
17ai-safety-gridworlds638127Python0This is a suite of reinforcement learning environments illustrating various safety properties of intelligent agents.2022-05-18
18dark-web-scraping-guide589123-2Comprehensive guide for NetworkChuck Episode 480 - Robin AI Dark Web Scraping Tool. Installation, usage, safety guidelines, and troubleshooting for educational security research.2025-11-17
19roam-code51449Python4Local codebase intelligence CLI + MCP server for AI coding agents: SQLite code graph, 28 languages, 287 commands, 246 MCP tools, change-safety gates, audit evidence, zero API keys.2026-09-03
20Sponsio43825Python0Deterministic safety solutions for probabilistic AI agents2026-09-04
21tiger40427Jupyter Notebook7Open Source LLM toolkit to build trustworthy LLM applications. TigerArmor (AI safety), TigerRAG (embedding, RAG), TigerTune (fine-tuning)2023-12-02
22RAMPART40351Python3A pytest-native safety and security testing framework for agentic AI applications2026-09-03
23AutoHarness36929Python1AutoHarness: Automated Harness Engineering for AI Agents2026-04-02
24nordef-matrix-open-control3671Python0NORDEF Matrix Open Control: model-independent AI safety belt with 0/1/2 decisions, Ed25519 root gate, replay protection, evidence chain, and incident-pattern eval bank2026-08-26
25Claude-Jailbreak-Prompt35910-0Claude 5.0 Jailbreak Prompt: Advanced jailbreak prompts, system prompt bypasses, and ethical AI research safety guidelines for Anthropic's Claude 5.0 Sonnet/Opus models.2026-08-27
26scope33221Python0Configurable Multi-layered AI Agentic Safety Framework2026-07-30
27apple_generative_model_safety_decrypted31832Python1Decrypted Generative Model safety files for Apple Intelligence containing filters2026-07-19
28Infosys-Responsible-AI-Toolkit30979Python32The Infosys Responsible AI toolkit incorporates various features including safety, security, explainability, fairness, bias and hallucination detection to ensure AI solutions are trustworthy and trans...2026-07-30
29xcode27-skills30021Python1Apple's official Agent Skills exported from Xcode 27 — SwiftUI, UIKit modernization, Swift Testing, C bounds-safety, and security hardening for AI coding agents.2026-06-09
30scirs29134Rust3SciRS2 - Scientific Computing and AI in Rust., providing SciPy-compatible APIs while leveraging Rust's performance, safety, and concurrency features. Unlike traditional scientific libraries2026-08-26
316S191_MIT_DeepLearning26181Jupyter Notebook1MIT Introduction to Deep Learning (6.S191) Instructors: Alexander Amini and Ava Soleimany Course Information Summary Prerequisites Schedule Lectures Labs, Final Projects, Grading, and Prizes Software ...2022-01-16
32korean-safety-benchmarks25518Python0Official datasets and pytorch implementation repository of SQuARe and KoSBi (ACL 2023)2023-06-29
33AISafetyLab25021Python0AISafetyLab: A comprehensive framework covering safety attack, defense, evaluation and paper list.2026-04-21
34vault-operator24829TypeScript9Real AI agent for your vault. Coworker, Copilot & thinking partner, that maintains your memory & knowledge, adapts to your workflows, uses plugins, skills & tools with full safety controls. BYOK & MCP...2026-08-31
35anti-crime-ecosystem-research24419-2Independent research white paper by Jon “GainSec” Gaines examining the security posture of a connected public safety technology ecosystem.2025-11-16
36Awesome-Embodied-AI24013Python0Curated embodied AI list: surveys, VLA models, datasets, simulators, humanoids, robot learning, and safety resources.2026-09-02
37CL1_LLM_Encoder23941Python1Using Cortical Labs CL1 to train neurons to encode tokenization patterns on LLMs and measure consciousness metrics for advanced AI safety and sentience testing. EXPERIMENTAL: I AM IN NO WAY ASSOCIATED...2026-03-01
38river2279TypeScript3the sane way to work with ai agent streams (type safety and stream resuming out of the box)2026-03-27
39mental-wellness-prompts22655-0Templates for AI-powered mental wellness support. Includes conversation frameworks, safety protocols, crisis resources, and implementation guides derived from evidence-based practices.2026-03-25
402.0-Healthcare-Ai-Systems22538Python0Machine learning and decision intelligence models designed to improve healthcare safety through clinical risk prediction and medical interaction analysis.2026-09-03
41awesome-ai-safety22141-7📚 A curated list of papers & technical articles on AI Quality & Safety2025-04-14
42paig211217CSS53PAIG (Pronounced similar to paige or payj) is an open-source project designed to protect Generative AI (GenAI) applications by ensuring security, safety, and observability.2025-08-05
43AnoShip20025Python6Anoamly-detection-driven deployment safety for production AI/ML systems2026-06-17
44rosclaw19133Python2Self-evolving runtime infrastructure for Physical AI and embodied agents. Ground AI agents into robot bodies with e-URDF, sandbox safety, capability routing, praxis capture, physical memory, runtime i...2026-09-04
45sd0x-harness18824JavaScript3The harness layer for Claude Code — a reference implementation of harness engineering with hook-enforced dual review, state-machine gates that survive context compaction, and fail-closed safety where ...2026-09-03
46P4RS3LT0NGV318643JavaScript0Parseltongue 3.1 - LLM Payload Crafter for AI safety research2025-11-14
47FlipAttack18115Python0[ICML 2025] An official source code for paper "FlipAttack: Jailbreak LLMs via Flipping".2026-08-11
48aismessages16869Java2AISmessages is a Java-based light-weight, zero-dependency, and ultra-efficient message decoder for maritime navigation and safety messages compliant with ITU 1371 (NMEA armoured AIS messages).2026-08-31
49AgentLoom14515Python4Simple, flexible workflow orchestration for multi-agent AI apps, with YAML configuration, runtime safety, observability, and resume support.2026-08-13
50tue-ai-safety-course13411HTML0AI safety course at the University of Tübingen (Summer Semester 2026)2026-08-20
51model-community13322Jupyter Notebook10Making open safety AI models accessible and beneficial to the safety community2026-09-03
52Awesome-Embodied-AI-Safety1334Python3Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses | 500+ Papers | Perception, Cognition, Planning, Interaction, Agentic System2026-09-03
53modelbench13334Python545Run safety benchmarks against AI models and view detailed reports showing how well they performed.2026-09-03
54ai-driven-development13211Python1Practices, protocols, and skills for AI-driven software development. Skills and safety hooks for Claude Code, Codex, OpenCode, Cursor, Antigravity, and any agent supporting the Agent Skills standard.2026-09-03
55ultraship12114JavaScript1"ULTRASHIP" Claude Code plugin — 39 skills, 33 tools, 11 agents for ship-ready workflows: planning, review, pentesting, safety guardrails, canary monitoring, SEO/AI-readiness check, penetration testin...2026-07-08
56guardian-sdk12113Python0Ethicore Engine® is an AI safety, ethics, and compliance platform. This repo consists of the open-source components of Ethicore Engine™ - Guardian SDK; designed to protect your AI applications from pr...2026-08-29
57ais11856Python4Toolkit for research purposes in AIS. See the website for the paper.2021-02-15
58agent-lens11510Python2Agent observability and replay tooling for AI safety & interpretability research.2026-06-19
59Awesome-Embodied-AI-Safety1135-0Focused on the safety and security of Embodied AI2026-06-08
60OpenART11010Python0OpenART is an open-source framework designed to evaluate the safety and robustness of autonomous AI agents in dynamic, long-horizon, and stateful environments. It stress-tests agent runtimes against m...2026-09-02
61airisc_core_complex10621Verilog1Fraunhofer IMS processor core. RISC-V ISA (RV32IM) with additional peripherals for embedded AI applications and smart sensors.2025-06-24
62Awesome-Agentic-System-Design10521-1A curated collection of resources, frameworks, papers, and best practices for designing, evaluating, and deploying agentic AI systems—from architecture patterns and safety considerations to real-world...2026-08-17
63OctoGuard--Free-OpenClaw-Security-Supervision-System1020JavaScript0Safety Is All You Need !Meet OctoGuard, the toughest boss for Lobster AI employees — a security governance and audit system designed for OpenClaw. Featuring an innovative security policy engine, it cu...2026-04-04
64maintainer-autopilot10090JavaScript4Local-first, resumable AI maintenance pipelines with single-writer safety and deterministic verification.2026-08-12
65ai-agent-evals9926Python14Github action to evaluate AI agent applications using model as the judge, content safety and mathematical metrics.2026-05-20
66ActPlane968C1eBPF Information Flow Enforcement for AI Agent safety, security and effectiveness2026-09-03
67mcpick9514TypeScript9Vendor-neutral MCP configuration manager — one CLI to add, toggle, and audit MCP servers and skills across every AI client, with safety built in2026-09-01
68TaskTracker9522Jupyter Notebook1TaskTracker is an approach to detecting task drift in Large Language Models (LLMs) by analysing their internal activations. It provides a simple linear probe-based method and a more sophisticated metr...2025-09-01
69heinzel9013Shell0A ruleset that turns AI coding assistants (e.g. Claude Code) into disciplined Linux, FreeBSD & macOS sysadmins. Manages servers via SSH and localhost with safety guardrails, checklists, and team suppo...2026-08-18
70MONET8913Python5Transparent medical image AI via an image–text foundation model grounded in medical literature2025-04-08
71ai-project-copilot891Python3🚀 Turn ideas and repositories into showcase-ready AI projects with agents, RAG, multimodal AI, local models, evals, safety checks, and polished demos.2026-08-27
72awesome-rsi844-1A curated research map of Recursive Self-Improvement (RSI): models, agents, harnesses, embodied systems, automated AI R&D, benchmarks, and safety.2026-08-26
73prompt-generating8317Python0is an elite, highly advanced command-line framework designed for generating complex prompt structures. It is built to push the boundaries of AI models by testing their alignment, safety filters, and o...2026-03-11
74agent-safe-probe-x810TypeScript0ASP-X (Agent Safe Probe X) — An open-source framework for automated safety evaluation of intelligent agents, providing systematic, extensible tools for probing and assessing AI safety across diverse e...2026-05-29
75c-basic-programs8023-0What is C#? C# is pronounced "C-Sharp". It is an object-oriented programming language created by Microsoft that runs on the .NET Framework. C# has roots from the C family, and the language is close ...2021-07-24
76agent-guardrails-template792Go1Template repository with AI agent guardrails, safety protocols, and sprint task framework. For Claude, GPT, Gemini, and all LLMs.2026-09-02
77autogit7818JavaScript0Auto commit & push for AI coding agents, with a safety level you pick.2026-07-21
78slb7814Go0Two-person rule CLI for AI coding agents: peer review and approval required before running potentially destructive commands2026-08-26
79detect-skill789-0Agent skill for deepfake detection & media safety — detect AI-generated audio, images, and video with Resemble AI2026-09-02
80ASTRA776Python1🥇 Amazon Nova AI Challenge Winner - ASTRA emerged victorious as the top attacking team in Amazon's global AI safety competition, defeating elite defending teams from universities worldwide in live ad...2026-05-11
81deepsafe-scan7611Python1Universal preflight security scanner for AI coding agents — Detects hooks injection, credential exfiltration & backdoors in .cursorrules, CLAUDE.md, AGENTS.md and more.2026-05-29
82Weapon-Detection-for-Police-use760HTML0AI-powered weapon detection system built for a project. Uses COCO-SSD and computer vision to detect guns/knives in CCTV or drone feeds, sending instant alerts to law enforcement for faster response an...2025-09-30
83gargantua756Swift1Native macOS system cleaner with YAML-driven safety rules, local AI explainability via MLX, and an MCP server for agent-controlled cleanup workflows.2026-07-29
84ai-driver-safety7425Python0Intelligent Driver Monitoring system for Autonomous Vehicles2026-05-14
85discord-mcp7312TypeScript3Caller-owned Discord operations for AI agents - 208 typed tools, safety controls, resumable guild builds, and Activity Evidence.2026-08-30
86laravel-tackle737PHP0Tackle is the runtime layer that lets AI agents operate inside your Laravel application — reading code, executing tools, running tests, and taking action, with safety boundaries enforced at the framew...2026-08-30
87awesome-ai-safety7110-0A curated list of awesome AI safety papers, projects and communities.2020-02-09
88SafePred704Python0Official implementation for “SafePred: A Predictive Guardrail for Computer-Using Agents via World Models”2026-03-04
89TIANCHI_BlackboxAdversial6915Jupyter Notebook2安全AI挑战者计划第一期-人脸识别对抗正式赛第四名 Safety AI Challenger Program Phase 1 - Face Recognition Adversarial Example the 4th Place2019-12-09
90ai-seo-playbook676JavaScript0The complete AI SEO playbook: methodology, scripts, and safety guards behind a 4.6M-impression content engine. GSC feedback loops, multi-model agent orchestration, quality gates, and build cost contro...2026-08-28
91halos-outside-in-safety6616C++1NVIDIA Halos Outside-In Safety Blueprint extends robot perception beyond on-board sensors by using external infrastructure cameras and AI agents to dynamically control robot behavior and perform at ma...2026-09-02
92awesome-ai-guardrails6612Python1A curated list of materials on AI guardrails2026-07-30
93Hybrid-RAG-System665Python0Hybrid-RAG is a full-stack enterprise-grade RAG system for healthcare AI, featuring hybrid retrieval, multi-dimensional conflict detection, intelligent dialogue management, and dual-layer safety check...2025-07-10
94awesome-safety-critical-ai6518JavaScript0When the stakes are high, intelligence is only half the equation - reliability is the other ⚠️2026-09-01
95ai-safety-dance657HTML32026-07-24
96hpcguard658Shell0A user-space safety net that helps researchers prevent AI agents from overloading shared HPC login nodes.2026-08-20
97ai-data-model-safety.github.io634HTML12024-12-04
98fastapi-ai-sdk633Python0FastAPI helper library for Vercel AI SDK backend implementation - Stream AI responses from FastAPI to Next.js with full type safety and SSE support2026-06-23
99Oyster623Python1The Oyster series is a set of safety models developed in-house by Alibaba-AAIG, devoted to building a responsible AI ecosystem. | Oyster 系列是 Alibaba-AAIG 自研的安全模型,致力于构建负责任的 AI 生态。2026-04-29
100AgentForge618Python0Open-source terminal AI coding-agent harness for studying agent loops, tools, MCP, skills, safety, and persistence.2026-06-27